Papers with non-English-speaking users

1 papers
Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework (2026.eacl-srw)

Copied to clipboard

Challenge: Currently, most vision-language models are trained on English-centric data, limiting their usability for non-English-speaking users.
Approach: They reproduce and adapt LLaVA-Next methodology to create Polish VLMs . they use a fully automated pipeline for translating and filtering existing multimodal datasets based on Polish data for OCR and culturally specific tasks.
Outcome: The proposed model improves on a Polish-adapted model and shows higher quality captions in generative evaluations.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations